Papers with Viable cross-lingual transfer

1 papers
JW300: A Wide-Coverage Parallel Corpus for Low-Resource Languages (P19-1)

Copied to clipboard

Challenge: a shortage of parallel data in low-resource languages creates a bottleneck for cross-lingual transfer . a massive collection of parallel texts for over 300 diverse languages is our main contribution .
Approach: They propose a parallel corpus of over 300 languages with 100 thousand parallel sentences per language pair on average.
Outcome: The proposed dataset can be used to build cross-lingual word embeddings and multi-source part-of-speech projections.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations